PAV ontology: provenance, authoring and versioning

نویسندگان

  • Paolo Ciccarese
  • Stian Soiland-Reyes
  • Khalid Belhajjame
  • Alasdair J. G. Gray
  • Carole A. Goble
  • Tim Clark
چکیده

BACKGROUND Provenance is a critical ingredient for establishing trust of published scientific content. This is true whether we are considering a data set, a computational workflow, a peer-reviewed publication or a simple scientific claim with supportive evidence. Existing vocabularies such as Dublin Core Terms (DC Terms) and the W3C Provenance Ontology (PROV-O) are domain-independent and general-purpose and they allow and encourage for extensions to cover more specific needs. In particular, to track authoring and versioning information of web resources, PROV-O provides a basic methodology but not any specific classes and properties for identifying or distinguishing between the various roles assumed by agents manipulating digital artifacts, such as author, contributor and curator. RESULTS We present the Provenance, Authoring and Versioning ontology (PAV, namespace http://purl.org/pav/): a lightweight ontology for capturing "just enough" descriptions essential for tracking the provenance, authoring and versioning of web resources. We argue that such descriptions are essential for digital scientific content. PAV distinguishes between contributors, authors and curators of content and creators of representations in addition to the provenance of originating resources that have been accessed, transformed and consumed. We explore five projects (and communities) that have adopted PAV illustrating their usage through concrete examples. Moreover, we present mappings that show how PAV extends the W3C PROV-O ontology to support broader interoperability. METHOD The initial design of the PAV ontology was driven by requirements from the AlzSWAN project with further requirements incorporated later from other projects detailed in this paper. The authors strived to keep PAV lightweight and compact by including only those terms that have demonstrated to be pragmatically useful in existing applications, and by recommending terms from existing ontologies when plausible. DISCUSSION We analyze and compare PAV with related approaches, namely Provenance Vocabulary (PRV), DC Terms and BIBFRAME. We identify similarities and analyze differences between those vocabularies and PAV, outlining strengths and weaknesses of our proposed model. We specify SKOS mappings that align PAV with DC Terms. We conclude the paper with general remarks on the applicability of PAV.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

An Ontology-based Framework for Author-learning Content Interaction

This paper presents an ontology-based framework for capturing information about interaction between authors and learning contents. The approach is primarily focused on keeping track of changes made to a learning content as well as information about authors and tools, which take part in the interaction. We introduce the Author-Learning Content Interaction (ALCI) ontology, which gives a formal re...

متن کامل

AO: An Open Annotation Ontology for Science on the Web

We present the Annotation Ontology (AO), an open ontology in OWL for annotating scientific documents on the web. AO supports both human and algorithmic content annotation. It enables “stand-off” (separate) metadata anchored to specific positions in a web document by any one of several methods. In AO, the document may be annotated but is not required to be under update control of the annotator. ...

متن کامل

Semantic Wikis Versioning with BiFröST

One of the powerful and popular tools that are used to support Collaborative Knowledge Engineering are Semantic Wikis. They are easily accessible and provide ACL mechanisms, but they lack a good versioning mechanism. In this paper an extended version of such a mechanism is proposed. Besides elements that appear in every Wiki system, like simple changelog and place for discussion, it incorporate...

متن کامل

The Quad Economy of a Semantic Web Ontology Repository

Quad stores have a number of features that make them very attractive for Semantic Web applications. Quad stores use Named Graphs to give contextual information to RDF graphs. Developers can use these contexts to incorporate meta-data such as provenance and versioning. The OWL language specification provides the owl:imports construct as one of the key ways to reuse the ontologies on the Semantic...

متن کامل

Extended MKCOL for Web Distributed Authoring and Versioning (WebDAV)

This specification extends the Web Distributed Authoring and Versioning (WebDAV) MKCOL (Make Collection) method to allow collections of arbitrary resourcetype to be created and to allow properties to be set at the same time.

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره 4  شماره 

صفحات  -

تاریخ انتشار 2013